Papers with video description generation
Incorporating Semantic Attention in Video Description Generation (L18-1)
Copied to clipboard
| Challenge: | Existing methods to generate video descriptions fail to mention objects and actions in videos . |
| Approach: | They propose an LSTM-based sequence-to-sequence model with semantic attention mechanism for video description generation that includes external fine-grained visual information. |
| Outcome: | The proposed model can selectively focus on external fine-grained visual information and have a better quality of video descriptions. |